Search CORE

1 research outputs found

Model-enhanced Vector Index

Author: Chang Ruiheng
Chen Qi
Cui Bin
Deng Weiwei
Ding Yang
Hou Yingyan
Miao Xupeng
Miao Ziming
Pang Bochen
Sun Hao
Wang Haonan
Wang Yujing
Xie Xing
Yang Fan
Yang Mao
Zhan Yuefeng
Zhang Hailin
Zhang Qi
Zhang Ting
Publication venue
Publication date: 23/09/2023
Field of study

Embedding-based retrieval methods construct vector indices to search for document representations that are most similar to the query representations. They are widely used in document retrieval due to low latency and decent recall performance. Recent research indicates that deep retrieval solutions offer better model quality, but are hindered by unacceptable serving latency and the inability to support document updates. In this paper, we aim to enhance the vector index with end-to-end deep generative models, leveraging the differentiable advantages of deep retrieval models while maintaining desirable serving efficiency. We propose Model-enhanced Vector Index (MEVI), a differentiable model-enhanced index empowered by a twin-tower representation model. MEVI leverages a Residual Quantization (RQ) codebook to bridge the sequence-to-sequence deep retrieval and embedding-based models. To substantially reduce the inference time, instead of decoding the unique document ids in long sequential steps, we first generate some semantic virtual cluster ids of candidate documents in a small number of steps, and then leverage the well-adapted embedding vectors to further perform a fine-grained search for the relevant documents in the candidate virtual clusters. We empirically show that our model achieves better performance on the commonly used academic benchmarks MSMARCO Passage and Natural Questions, with comparable serving latency to dense retrieval solutions

arXiv.org e-Print Archive